Papers by Yash Mahesh Bangera

1 papers
Language Variety Identification with True Labels (2024.lrec-main)

Copied to clipboard

Challenge: Language identification datasets are compiled with the assumption that the gold label of each instance is determined by where texts are retrieved from.
Approach: They present a human-annotated multilingual dataset for language variety identification . they use a model to train multiple models to discriminate between different languages .
Outcome: The proposed dataset provides a reliable benchmark toward robust and fairer language variety identification systems.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations